Ensemble Voting System for Multiclass protein fold Recognition

نویسندگان

  • Yuehui Chen
  • Feng Chen
  • Jack Y. Yang
  • Mary Yang
چکیده

Protein structure classification is an important issue in understanding the associations between sequence and structure as well as possible functional and evolutionary relationships. Recently structural genomes initiatives and other high-throughput experiments have populated the biological databases at a rapid pace. In this paper, three types of classifiers, k nearest neighbors, class center and nearest neighbor and probabilistic neural networks and their homogenous ensemble for multiclass protein fold recognition problem are evaluated firstly, and then a heterogenous ensemble Voting System is designed for the same problem. The different features and/or their combinations extracted from the protein fold dataset are used in these classification models. The heterogenous classification results are then put into a voting system to get the final result. The experimental results show that the proposed method can improve prediction accuracy by 4%–10% on a benchmark dataset containing 27 SCOP folds.

برای دانلود متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

منابع مشابه

Enhancing Protein Fold Prediction Accuracy Using an Ensemble of Different Classifiers

Protein fold prediction problem is considered as a key point to protein structure recognition and structural discoveries. Recent advances in pattern recognition field brought a great interest to apply pattern classification techniques to tackle this problem. From the pattern recognition point of view, the protein fold prediction problem can be expressed as a multi-class classification task that...

متن کامل

Application of ensemble learning techniques to model the atmospheric concentration of SO2

In view of pollution prediction modeling, the study adopts homogenous (random forest, bagging, and additive regression) and heterogeneous (voting) ensemble classifiers to predict the atmospheric concentration of Sulphur dioxide. For model validation, results were compared against widely known single base classifiers such as support vector machine, multilayer perceptron, linear regression and re...

متن کامل

Multi-class protein fold recognition using support vector machines and neural networks

MOTIVATION Protein fold recognition is an important approach to structure discovery without relying on sequence similarity. We study this approach with new multi-class classification methods and examined many issues important for a practical recognition system. RESULTS Most current discriminative methods for protein fold prediction use the one-against-others method, which has the well-known '...

متن کامل

Multiclass Vehicle Type Recognition System using an Oriented Points Set based Model

This paper presents a framework for multi-class vehicle type identification based on oriented contour points. The decision process combines four classifiers: three voting algorithms and a distance error. This method have been tested on a realistic data set obtaining similar results for equivalent recognition frameworks with different features selections [11]. A confidence criterion is also calc...

متن کامل

Ensemble neural classifier design for face recognition

A method for tuning MLP learning parameters in an ensemble classifier framework is presented. No validation set or cross-validation technique is required to optimize parameters for generalisability. In this paper, the technique is applied to face recognition using Error-Correcting Output Coding strategy to solve multiclass problems.

متن کامل

ذخیره در منابع من


  با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید

برای دانلود متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

عنوان ژورنال:
  • IJPRAI

دوره 22  شماره 

صفحات  -

تاریخ انتشار 2008